Embedded Implementation of a Hybrid Neural-network Telephone Speech Recognition System

نویسندگان

  • Johan Schalkwyk
  • Pieter Vermeulen
  • Mark Fanty
چکیده

TELEPHONE SPEECH RECOGNITION SYSTEM Johan Schalkwyk, Pieter Vermeulen, Mark Fanty and Ronald Coley Center for Spoken Language Understanding, Oregon Graduate Institute of Science and Technology 20000 N.W. Walker Road, P.O. Box 91000, Portland, OR 97291-1000, USA Tel: +1 503-6901318, E-mail: [email protected] ABSTRACT An embedded implementation of a hybrid neural-network based telephone speech recognition system is described. The system (1) digitizes speech and computes a PLP spectral representation; (2) uses a Neural Network to assign 534 contextual phoneme probabilities; and (3) uses these phoneme probabilities for a general purpose word-spotting task.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Persian Phone Recognition Using Acoustic Landmarks and Neural Network-based variability compensation methods

Speech recognition is a subfield of artificial intelligence that develops technologies to convert speech utterance into transcription. So far, various methods such as hidden Markov models and artificial neural networks have been used to develop speech recognition systems. In most of these systems, the speech signal frames are processed uniformly, while the information is not evenly distributed ...

متن کامل

بهبود عملکرد سیستم بازشناسی گفتار پیوسته بوسیله ویژگی‌های استخراج شده از مانیفولدهای گفتاری در فضای بازسازی شده فاز

The design for new feature extraction methods out of the speech signal and combination of their obtained information is one of the most effective approaches to improve the performance of automatic speech recognition (ASR) system. Recent researches have been shown that the speech signal contains nonlinear and chaotic properties, but the effects of these properties are not used in the continuous ...

متن کامل

شبکه عصبی پیچشی با پنجره‌های قابل تطبیق برای بازشناسی گفتار

Although, speech recognition systems are widely used and their accuracies are continuously increased, there is a considerable performance gap between their accuracies and human recognition ability. This is partially due to high speaker variations in speech signal. Deep neural networks are among the best tools for acoustic modeling. Recently, using hybrid deep neural network and hidden Markov mo...

متن کامل

A Comparison of Hmm and Neural Network Approaches to Real World Telephone Speech Applications

WORLD TELEPHONE SPEECH APPLICATIONS Pieter Vermeulen, Etienne Barnard, Yonghong Yan, Mark Fanty and Ronald Coley Center for Spoken Language Understanding, Oregon Graduate Institute of Science and Technology 20000 N.W. Walker Road, P.O. Box 91000, Portland, OR 97291-1000, USA Tel: +1 503-6901484, E-mail: [email protected] ABSTRACT We compare a standard HMM based and a neural network based appro...

متن کامل

The Implementation of Speech Recognition Systems on Fpga-based Embedded Systems with Soc Architecture

An implementation of Artificial-Neural-Network (ANN) based speech recognition systems on the embedded platform is explored in this paper. An FPGA chip is adopted as the hardware of the embedded platform with the architecture of SOC. This makes the speech recognition systems applicable on the voice activated systems in toys, games, smart phones, office devices, vehicular communications, etc. Bec...

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره   شماره 

صفحات  -

تاریخ انتشار 1995